Tandem-repeat protein domains across the tree of life
نویسندگان
چکیده
Tandem-repeat protein domains, composed of repeated units of conserved stretches of 20-40 amino acids, are required for a wide array of biological functions. Despite their diverse and fundamental functions, there has been no comprehensive assessment of their taxonomic distribution, incidence, and associations with organismal lifestyle and phylogeny. In this study, we assess for the first time the abundance of armadillo (ARM) and tetratricopeptide (TPR) repeat domains across all three domains in the tree of life and compare the results to our previous analysis on ankyrin (ANK) repeat domains in this journal. All eukaryotes and a majority of the bacterial and archaeal genomes analyzed have a minimum of one TPR and ARM repeat. In eukaryotes, the fraction of ARM-containing proteins is approximately double that of TPR and ANK-containing proteins, whereas bacteria and archaea are enriched in TPR-containing proteins relative to ARM- and ANK-containing proteins. We show in bacteria that phylogenetic history, rather than lifestyle or pathogenicity, is a predictor of TPR repeat domain abundance, while neither phylogenetic history nor lifestyle predicts ARM repeat domain abundance. Surprisingly, pathogenic bacteria were not enriched in TPR-containing proteins, which have been associated within virulence factors in certain species. Taken together, this comparative analysis provides a newly appreciated view of the prevalence and diversity of multiple types of tandem-repeat protein domains across the tree of life. A central finding of this analysis is that tandem repeat domain-containing proteins are prevalent not just in eukaryotes, but also in bacterial and archaeal species.
منابع مشابه
Data Mining for Identification of Forkhead Box O (FOXO3a) in Different Organisms Using Nucleotide and Tandem Repeat Sequences
Background: Deregulation of FOXO3a gene which belongs to Forkhead box O (FOXO) transcription factors, can cause cancer (e.g. breast cancer). FOXO factors have important role in ubiquitination, acetylation, de-acetylation, protein-protein interactions and phosphorylation. Understanding the regulation and mechanisms of FOXO3a can lead to cancer treatment. The aim of this study recent association...
متن کاملExpansion and Function of Repeat Domain Proteins During Stress and Development in Plants
The recurrent repeats having conserved stretches of amino acids exists across all domains of life. Subsequent repetition of single sequence motif and the number and length of the minimal repeating motifs are essential characteristics innate to these proteins. The proteins with tandem peptide repeats are essential for providing surface to mediate protein-protein interactions for fundamental biol...
متن کاملMore than one tandem repeat domain of Extracellular Adherence Protein (Eap) of Staphylococcus aureus is required for aggregation, adherence and host cell invasion, but not for leucocyte activation
Eap is a multifunctional Staphylococcus aureus protein and broad spectrum adhesin for several host matrix and plasma proteins. We investigated the interactions of full length Eap and five recombinant tandem repeat domains with host proteins using surface plasmon resonance (SPR; BIAcore®) and ligand overlay assays. In addition, agglutination and host cell interaction, namely 5 adherence, invasio...
متن کاملGenetic Variation of Informative Short Tandem Repeat (STR) Loci in an Iranian Population
In the present study, genotyping of six short tandem repeat (STR) loci including CSF1PO, D16S539, F13A01, F13B, LPL and HPRTB was performed on genomic DNA from 127 unrelated individuals from the Iranian province of Isfahan. The results indicated that the allele and genotype distributions were in accordance with Hardy-Weinberg expectations. The observed heterozygosity (Ho), expected heterozygosi...
متن کاملVNTR9 and VNTR10, two newly-found variable-number tandem repeat loci useful in MLVA genotyping of Bordetella pertussis
Background & Aims: Bordetella pertussis, the causative agent of whooping cough, continues to infect human hosts even in those populations where infants and children are routinely vaccinated. Causes of pertussis epidemiology are not fully identified unless strains of the pathogen are characterized by molecular means. Golbally, Multi Locus Variable Number of Tandem Repeats analysis (MLVA) has pro...
متن کامل